Papers with sanity checking models
Make Up Your Mind! Adversarial Generation of Inconsistent Natural Language Explanations (2020.acl-main)
Copied to clipboard
| Challenge: | a promising research direction consists of designing neural models capable of generating natural language explanations for their predictions. |
| Approach: | They propose a framework for sanity checking models against inconsistent explanations . they apply the framework to a state-of-the-art neural natural language inference model . |
| Outcome: | The proposed framework can generate inconsistent explanations on a state-of-the-art model . it also addresses the problem of adversarial attacks with full target sequences . |